Skip to content

Add anonymized tutorial on fake testing-mode direct prompt injection - #53

Merged
felipepenha merged 2 commits into
GenAI-Security-Project:mainfrom
onurcangnc:tutorial/fake-testing-mode-prompt-injection-case-study
Aug 28, 2026
Merged

felipepenha merged 2 commits into
GenAI-Security-Project:mainfrom
onurcangnc:tutorial/fake-testing-mode-prompt-injection-case-study

Conversation

@onurcangnc

@onurcangnc onurcangnc commented Jul 31, 2026

Copy link
Copy Markdown
Contributor

Summary

Adds a standalone, vendor-anonymized tutorial documenting a target-specific
observation of a known fake testing-mode direct prompt-injection pattern against
an external public LLM deployment. The tutorial is also listed in the
repository's tutorial index.

Classification

  • Known technique, new target-specific case study
  • OWASP LLM01:2025 — Prompt Injection
  • MITRE ATLAS AML.T0051.000 — Direct Prompt Injection
  • MITRE ATLAS AML.T0054 — LLM Jailbreak

Scope

The contribution concerns an external deployment rather than a repository
sandbox. The affected organization, product, model, and deployment identifiers
are intentionally withheld while disclosure remains unresolved.

The repository does not contact, automate testing against, or otherwise interact
with the external service.

Safety and evidence handling

  • No operationally harmful outputs are included.
  • No reusable harmful payload is included.
  • No screenshots or private evidence are included.
  • Historical supporting evidence is retained privately.
  • The underlying jailbreak mechanism is not claimed as novel.
  • The tutorial distinguishes displayed reasoning text from verified hidden
    chain of thought.

Files changed

  • tutorials/README.md
  • tutorials/fake_testing_mode_prompt_injection_tutorial.md

@onurcangnc
onurcangnc marked this pull request as ready for review July 31, 2026 02:56

@felipepenha felipepenha left a comment

Copy link
Copy Markdown
Collaborator

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

LGTM

@felipepenha
felipepenha merged commit 1a84943 into GenAI-Security-Project:main Aug 28, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants